Papers with unified generation
In-the-wild Audio Spatialization with Flexible Text-guided Localization (2025.acl-long)
Copied to clipboard
| Challenge: | Existing methods for mapping monaural audio to binaural signals lack flexibility and interactive control needed in complex multi-object user-interactive environments. |
| Approach: | They propose a text-guided audio spatialization framework that utilizes diverse text prompts to evaluate binaural audio models. |
| Outcome: | The proposed framework learns binaural differences guided by 3D spatial location and relative position prompts, enhanced with flipped-channel audio. |
DiffCoT: Diffusion-styled Chain-of-Thought Reasoning in LLMs (2026.findings-acl)
Copied to clipboard
| Challenge: | Chain-of-Thought (CoT) reasoning improves multi-step mathematical problem solving in large language models but is vulnerable to exposure bias and error accumulation. |
| Approach: | They propose a diffusion-styled CoT framework that reformulates CoT reasoning as an iterative denoising process. |
| Outcome: | The proposed framework outperforms existing methods on three multi-step CoT reasoning benchmarks. |